Tag
12 articles
Learn how to use AI coding agents to modernize research software while understanding the critical importance of human verification for scientific accuracy.
Supabase has open-sourced supabase/evals, a new benchmark that evaluates AI coding agents like Claude Code, Codex, and OpenCode using real Supabase tasks in containerized environments.
OpenAI's new field report demonstrates that coding agents, including Codex and Claude Code, can significantly reduce development runtimes in scientific computing projects.
Learn how to use AI coding agents for scientific computing by creating data analysis scripts, visualizations, and automated workflows using Python and common scientific libraries.
Researchers from Nvidia, Carnegie Mellon University, and UC Berkeley have developed a method for robots to train themselves using AI coding agents, achieving up to 99% success on complex grasping tasks.
This article explains how AI coding agents can find the right file but often miss the exact lines of code that need fixing, showing the limitations of current AI tools in understanding code context.
AI coding agents are reshaping software development in 2026, with platforms like Atoms, Devin, Windsurf, Cursor, and Warp leading the charge. These tools automate complex tasks and enhance developer productivity.
AI coding agent Devin, created by Cognition, is praised for its capabilities, but founder Scott Wu argues it shouldn't replace human programmers. Human expertise remains essential for complex decision-making and strategic planning in software development.
Cognition, the maker of AI coding agent Devin, has raised over $1 billion, pushing its valuation to $26 billion in under nine months. The funding reflects strong investor interest in AI development tools, though the real-world utility of such agents remains debated.
Programmer George Hotz warns that AI coding agents will become one of the industry's most costly mistakes, citing their inability to handle code details reliably. His critique highlights the growing divide in the tech community over AI's role in software development.
OpenAI has been named a Leader in the 2026 Gartner Magic Quadrant for Enterprise AI Coding Agents, with Codex recognized for innovation and enterprise-scale deployment.
This article explains spec-driven development, a methodology where formal specifications guide AI coding agents to generate reliable, production-ready code. It explores how tools like AWS Kiro and GSD are transforming software development in 2026.